Papers with Automatic text generation
Can Synthetic Text Help Clinical Named Entity Recognition? A Study of Electronic Health Records in French (2023.eacl-main)
Copied to clipboard
| Challenge: | In sensitive domains, the sharing of corpora is restricted due to confidentiality, copyrights or trade secrets. |
| Approach: | They use auto-regressive neural models to generate a clinical case corpus annotated with clinical entities and evaluate it for a named entity recognition task. |
| Outcome: | The proposed model can produce clinical case corpus annotated with clinical entities while maintaining confidentiality. |
A Benchmark Corpus for the Detection of Automatically Generated Text in Academic Publications (2022.lrec-1)
Copied to clipboard
| Challenge: | Automated text generation has achieved performance levels that make the generated text almost indistinguishable from those written by humans. |
| Approach: | They propose to use a completely synthetic dataset and a partial text substitution dataset to evaluate the quality of the generated research content. |
| Outcome: | The proposed datasets compare the generated texts to aligned original texts using fluency metrics such as BLEU and ROUGE. |